OcrV1, Main, Exploration, bibRecord, 001540

Dynamic Character Model Generation for Document Keyword Spotting

Identifieur interne : 001540 ( Main/Exploration ); précédent : 001539; suivant : 001541

Dynamic Character Model Generation for Document Keyword Spotting

Auteurs : Beom-Joon Cho [Corée du Sud] ; Bong-Kee Sin [Corée du Sud]

Source :

Lecture Notes in Computer Science [ 0302-9743 ] ; 2004.

RBID : ISTEX:B4083CD9ED82B414FA477B425C305B2043D58535

Abstract

Abstract: This paper proposes a novel method of generating statistical Korean Hangul character models in real time. From a set of grapheme average images we compose any character images, and then convert them to P2DHMMs. The nonlinear, 2D composition of letter models in Hangul is not straightforward and has not been tried for machine-print character recognition. It is obvious that the proposed method of character modeling is more advantageous than whole character or word HMMs in regard to the memory requirement as well as the training difficulty. In the proposed method individual character models are synthesized in real-time using the trained grapheme image templates. The proposed method has been applied to key character/word spotting in document images. In a series of preliminary experiments, we observed the performance of 86% and 84% in single and multiple word spotting respectively without language models. This performance, we believe, is adequate and the proposed method is effective for the real time keyword spotting applications

Url:

https://api.istex.fr/document/B4083CD9ED82B414FA477B425C305B2043D58535/fulltext/pdf

DOI: 10.1007/978-3-540-27868-9_123

Affiliations:

Corée du Sud

Links toward previous steps (curation, corpus...)

to stream Istex, to step Corpus: 000776
to stream Istex, to step Curation: 000767
to stream Istex, to step Checkpoint: 000D75
to stream Main, to step Merge: 001591
to stream Main, to step Curation: 001540

Le document en format XML

<record><TEI wicri:istexFullTextTei="biblStruct"><teiHeader><fileDesc><titleStmt><title xml:lang="en">Dynamic Character Model Generation for Document Keyword Spotting</title>
<author><name sortKey="Cho, Beom Joon" sort="Cho, Beom Joon" uniqKey="Cho B" first="Beom-Joon" last="Cho">Beom-Joon Cho</name>
</author>
<author><name sortKey="Sin, Bong Kee" sort="Sin, Bong Kee" uniqKey="Sin B" first="Bong-Kee" last="Sin">Bong-Kee Sin</name>
</author>
</titleStmt>
<publicationStmt><idno type="wicri:source">ISTEX</idno>
<idno type="RBID">ISTEX:B4083CD9ED82B414FA477B425C305B2043D58535</idno>
<date when="2004" year="2004">2004</date>
<idno type="doi">10.1007/978-3-540-27868-9_123</idno>
<idno type="url">https://api.istex.fr/document/B4083CD9ED82B414FA477B425C305B2043D58535/fulltext/pdf</idno>
<idno type="wicri:Area/Istex/Corpus">000776</idno>
<idno type="wicri:Area/Istex/Curation">000767</idno>
<idno type="wicri:Area/Istex/Checkpoint">000D75</idno>
<idno type="wicri:doubleKey">0302-9743:2004:Cho B:dynamic:character:model</idno>
<idno type="wicri:Area/Main/Merge">001591</idno>
<idno type="wicri:Area/Main/Curation">001540</idno>
<idno type="wicri:Area/Main/Exploration">001540</idno>
</publicationStmt>
<sourceDesc><biblStruct><analytic><title level="a" type="main" xml:lang="en">Dynamic Character Model Generation for Document Keyword Spotting</title>
<author><name sortKey="Cho, Beom Joon" sort="Cho, Beom Joon" uniqKey="Cho B" first="Beom-Joon" last="Cho">Beom-Joon Cho</name>
<affiliation wicri:level="1"><country xml:lang="fr">Corée du Sud</country>
<wicri:regionArea>Department of Computer Engineering, Chosun University, 501-759, Dong-ku, Gwangju</wicri:regionArea>
<wicri:noRegion>Gwangju</wicri:noRegion>
</affiliation>
<affiliation wicri:level="1"><country wicri:rule="url">Corée du Sud</country>
</affiliation>
</author>
<author><name sortKey="Sin, Bong Kee" sort="Sin, Bong Kee" uniqKey="Sin B" first="Bong-Kee" last="Sin">Bong-Kee Sin</name>
<affiliation wicri:level="1"><country xml:lang="fr">Corée du Sud</country>
<wicri:regionArea>Department of Computer Multimedia, Pukyong National University, Dayon-dong 599-1, Nam-ku, 608-737, Busan</wicri:regionArea>
<wicri:noRegion>Busan</wicri:noRegion>
</affiliation>
<affiliation wicri:level="1"><country wicri:rule="url">Corée du Sud</country>
</affiliation>
</author>
</analytic>
<monogr></monogr>
<series><title level="s">Lecture Notes in Computer Science</title>
<imprint><date>2004</date>
</imprint>
<idno type="ISSN">0302-9743</idno>
<idno type="eISSN">1611-3349</idno>
<idno type="ISSN">0302-9743</idno>
</series>
<idno type="istex">B4083CD9ED82B414FA477B425C305B2043D58535</idno>
<idno type="DOI">10.1007/978-3-540-27868-9_123</idno>
<idno type="ChapterID">123</idno>
<idno type="ChapterID">Chap123</idno>
</biblStruct>
</sourceDesc>
<seriesStmt><idno type="ISSN">0302-9743</idno>
</seriesStmt>
</fileDesc>
<profileDesc><textClass></textClass>
<langUsage><language ident="en">en</language>
</langUsage>
</profileDesc>
</teiHeader>
<front><div type="abstract" xml:lang="en">Abstract: This paper proposes a novel method of generating statistical Korean Hangul character models in real time. From a set of grapheme average images we compose any character images, and then convert them to P2DHMMs. The nonlinear, 2D composition of letter models in Hangul is not straightforward and has not been tried for machine-print character recognition. It is obvious that the proposed method of character modeling is more advantageous than whole character or word HMMs in regard to the memory requirement as well as the training difficulty. In the proposed method individual character models are synthesized in real-time using the trained grapheme image templates. The proposed method has been applied to key character/word spotting in document images. In a series of preliminary experiments, we observed the performance of 86% and 84% in single and multiple word spotting respectively without language models. This performance, we believe, is adequate and the proposed method is effective for the real time keyword spotting applications</div>
</front>
</TEI>
<affiliations><list><country><li>Corée du Sud</li>
</country>
</list>
<tree><country name="Corée du Sud"><noRegion><name sortKey="Cho, Beom Joon" sort="Cho, Beom Joon" uniqKey="Cho B" first="Beom-Joon" last="Cho">Beom-Joon Cho</name>
</noRegion>
<name sortKey="Cho, Beom Joon" sort="Cho, Beom Joon" uniqKey="Cho B" first="Beom-Joon" last="Cho">Beom-Joon Cho</name>
<name sortKey="Sin, Bong Kee" sort="Sin, Bong Kee" uniqKey="Sin B" first="Bong-Kee" last="Sin">Bong-Kee Sin</name>
<name sortKey="Sin, Bong Kee" sort="Sin, Bong Kee" uniqKey="Sin B" first="Bong-Kee" last="Sin">Bong-Kee Sin</name>
</country>
</tree>
</affiliations>
</record>

Pour manipuler ce document sous Unix (Dilib)

EXPLOR_STEP=$WICRI_ROOT/Ticri/CIDE/explor/OcrV1/Data/Main/Exploration

HfdSelect -h $EXPLOR_STEP/biblio.hfd -nk 001540 | SxmlIndent | more

HfdSelect -h $EXPLOR_AREA/Data/Main/Exploration/biblio.hfd -nk 001540 | SxmlIndent | more

Pour mettre un lien sur cette page dans le réseau Wicri

{{Explor lien
   |wiki=    Ticri/CIDE
   |area=    OcrV1
   |flux=    Main
   |étape=   Exploration
   |type=    RBID
   |clé=     ISTEX:B4083CD9ED82B414FA477B425C305B2043D58535
   |texte=   Dynamic Character Model Generation for Document Keyword Spotting
}}

This area was generated with Dilib version V0.6.32.
Data generation: Sat Nov 11 16:53:45 2017. Site generation: Mon Mar 11 23:15:16 2024

	Serveur d'exploration sur l'OCR
	Attention, ce site est en cours de développement ! Attention, site généré par des moyens informatiques à partir de corpus bruts. Les informations ne sont donc pas validées.

Serveur d'exploration sur l'OCR

Dynamic Character Model Generation for Document Keyword Spotting

Dynamic Character Model Generation for Document Keyword Spotting

Source :

Abstract

Links toward previous steps (curation, corpus...)

Le document en format XML

Pour manipuler ce document sous Unix (Dilib)

Pour mettre un lien sur cette page dans le réseau Wicri